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Abstract 

We consider the Erdos-Renyi random graph G{n,p) inside the critical window, where p = 
l/n + Xn.-'^^'-' for some A G R. We proved in [1] that considering the connected components of 
G{n,p) as a sequence of metric spaces with the graph distance rescaled by n~^^^ and letting 
n — > oo yields a non-trivial sequence of limit metric spaces C — (Ci ,€2, ■ ■ ■)■ These limit metric 
spaces can be constructed from certain random real trees with vertex-identifications. For a single 
such metric space, we give here two equivalent constructions, both of which are in terms of more 
standard probabilistic objects. The first is a global construction using Dirichlet random variables 
and Aldous' Brownian continuum random tree. The second is a recursive construction from an 
inhomogeneous Poisson point process on R+. These constructions allow us to characterize the 
distributions of the masses and lengths in the constituent parts of a limit component when it 
is decomposed according to its cycle structure. In particular, this strengthens results of Luczak 
et al. [29] by providing precise distributional convergence for the lengths of paths between kernel 
vertices and the length of a shortest cycle, within any fixed limit component. 

1 Introduction 

The Erdos-Renyi random graph G{n,p) is the random graph on vertex set {1,2, . . . , n} in which 
each of the (2) possible edges is present independently of the others with probability p. In the 50 
years since its introduction [20], this simple model has given rise to a very rich body of mathematics. 
(See the books [14, 26] for a small sample of this corpus.) In a previous paper [1], we considered the 
rescaled global structure of G{n,p) for p in the critical window - that is, where p = 1/n + Xn^^^^ for 
some A € M - when individual components are viewed as metric spaces with the usual graph distance. 
(See [1] for a discussion of the significance of the random graph phase transition and the critical 
window.) The subject of the present paper is the asymptotic behavior of individual components of 
G{n,p), again viewed as metric spaces, when p is in the critical window. 

Let C", C2 , . . . be the connected components of G{n, p) listed in decreasing order of size, with ties 
broken arbitrarily. Write C" = (Cf ,C^, . . .) and write n^i/^C" to mean the sequence of components 
viewed as metric spaces with the graph distance in each multiplied by n~^^^. Let dan be the 
Gromov-Hausdorff distance between two compact metric spaces (see [1] for a definition) . 

Theorem 1 ([1]). There exists a random sequence C of compact metric spaces such that as n ^ 00, 

^-l/3^n 4 
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where the convergence is in distribution in the distance d specified by 
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d{A,B)= [Y,dGH{A,B,f 



We refer to the individual metric spaces in the sequence C as the components of C. The proof of 
Theorem 1 reUes on a decomposition of any connected labeled graph G into two parts: a "canonical" 
spanning tree (see [1] for a precise definition of this tree), and a collection of additional edges which 
we call surplus edges. Correspondingly, the limiting sequence of metric spaces has a surprisingly 
simple description as a collection of random real trees (given below) in which certain pairs of vertices 
have been identified (vertex-identification being the natural analog of adding a surplus edge, since 
edge-lengths are converging to in the limit). 

In this paper, we consider the structure of the individual components of the limit C in greater 
detail. In the limit, these components have a scaling property which means that, in order to describe 
the distributional structure of a component, only the number of vertex identifications (which we 
also call the surplus) matters, and not the total mass of the tree in which the identifications take 
place. The major contribution of this paper is the description and justification of two construction 
procedures for building the components of C directly, conditional on their size and surplus. The 
importance of these new procedures is that instead of relying on a decomposition of a component 
into a spanning tree and surplus, they rely on a decomposition according to the cycle structure, 
which from many points of view is much more natural. 

The procedure we describe first is based on glueing randomly rescaled Brownian CRT's along 
the edges of a random kernel (see Section 2.2 for the definition of a kernel). This procedure is more 
combinatorial, and implicitly underlying it is a novel finite construction of a component of G{n,p), 
by first choosing a random kernel and then random doubly-rooted trees which replace the edges of 
the kernel. (However, we do not spell out the details of the finite construction since it does not lead 
to any further results.) It is this procedure that yields the strengthening of the results of Luczak, 
Pittel, and Wierman [29] . 

The second procedure contains Aldous' stick-breaking inhomogcncous Poisson process construc- 
tion of the Brownian CRT as a special case. Aldous' construction, first described in [2], has seen 
numerous extensions and applications, among which the papers of Aldous [5], Aldous, Miermont, 
and Pitman [8], Peres and Revelle [30], Schweinsberg [34] are notable. In particular, in the same 
way that the Brownian CRT arises as the limit of the uniform spanning tree in of Z'^ for c? > 4 
(proved in [34]), we expect our generalization to arise as the scaling limit of the components of 
critical percolation in or the d-dimensional torus, for large d. 

Before we move on to the precise description of the constructions, we introduce them informally 
and discuss their relationship with various facts about random graphs and the Brownian continuum 
random tree. 

1.1 Overview of the results 

A key object in this paper is Aldous' Brownian continuum random tree (CRT) [2-4]. In Section 2, 
we will give a full definition of the Brownian CRT in the context of real trees coded by excursions. 
For the moment, however, we will simply note that the Brownian CRT is encoded by a standard 
Brownian excursion, and give a more readily understood definition using a construction given in [3] . 

Stick-breaking construction of the Brownian CRT. Consider an inhomogeneous Poisson 
process on M"*" with instantaneous rate t at t £ M+. Let Ji, J2, . . . be its inter-jump times, in the 
order they occur ( Ji being measured from 0). Now construct a tree as follows. First take a (closed) 
line-segment of length Ji . Then attach another line-segment of length J2 to a uniform position on 
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the first line-segment. Attacli subsequent line-segments at uniform positions on the whole of the 
structure already created. Finally, take the closure of the object obtained. 

The canonical spanning tree appearing in the definition of a component of C is not the Brownian 
CRT, except when the surplus is 0; in general, its distribution is defined instead as a modification 
(via a change of measure) of the distribution of the Brownian CRT which favors trees encoded 
by excursions with a large area (see Section 2 for details). We refer to such a continuum random 
tree as a tilted tree. One of the main points of the present work is to establish strong similarity 
relationships between tilted trees and the Brownian CRT which go far beyond the change of measure 
in the definition. 

Our first construction focuses on a combinatorial decomposition of a connected graph into its 
cycle structure (kernel) and the collection of trees obtained by breaking down the component at 
the vertices of the kernel. In the case of interest here, the trees are randomly rescaled instances of 
Aldous' Brownian CRT. This is the first direct link between the components of C having strictly 
positive surplus and the Brownian CRT. 

As we already mentioned, our second construction extends the stick- breaking construction of the 
Brownian CRT given above. We prove that the tilted tree corresponding to a connected component 
of C with a fixed number of surplus edges can be built in a similar way: the difference consists in a 
bias in the lengths of the first few intervals in the process, but the rate of the Poisson point process 
used to split the remainder of M"*" remains unchanged. It is important to note that an arbitrary 
bias in the first lengths does not, in general, give a consistent construction: if the distribution is 
not exactly right, then the initial, biased lengths will appear too short or too long in comparison to 
the intervals created by the Poisson process. Indeed, it was not a priori obvious to the authors that 
such a distribution must necessarily exist. The consistency of this construction is far from obvious 
and a fair part of this paper is devoted to proving it. In particular, the results we obtain for the 
combinatorial construction (kernel/trees) identify the correct distributions for the first lengths. Note 
in passing that the construction shows that the change of measure in the definition of tilted trees is 
entirely accounted for by biasing a (random) number of paths in the tree. 

The first few lengths mentioned above are the distances between vertices of the kernel of the 
component. One can then see the stick-breaking construction as jointly building the trees: each 
interval chooses a partially-formed tree with probability proportional to the sum of its lengths (the 
current mass), and then chooses a uniformly random point of attachment in that tree. We show that 
one can analyze this procedure precisely via a continuous urn process where the bins are the partial 
trees, each starting initially with one of the first lengths mentioned above. The biased distribution of 
the initial lengths in the bins ensures that the process builds trees in the combinatorial construction 
which are not only Brownian CRT's but also have the correct joint distribution of masses. The proof 
relies on a decoupling argument related to de Finetti's theorem [9, 16]. 

1.2 Plan of the paper 

The two constructions we have just informally introduced are discussed precisely in Section 2. Along 
the way. Section 2 also introduces many of the key concepts and definitions of the paper. The 
distributional results are stated in Section 3. The remainder of the document is devoted to proofs. 
In Section 4 we derive the distributions of the lengths in a component of C between vertices of the 
kernel. The stick-breaking construction of a component of C is then justified in Section 5. Finally, 
our results about the distributions of masses and lengths of the collection of trees obtained when 
breaking down the cycle structure at its vertices are proved in Section 6. 
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2 Two constructions 



Suppose that Gf^ is a (connected) component of G{n,p) conditioned to have size (number of vertices) 
m < n. Theorem 1 entails that Gf„, with mn"^^^ — > a and pn — 1 as n — > cx) and distances rescaled 
by n~^/^, converges in distribution to some hmiting metric space, in the Gromov-HausdorfF sense. 
(We shall see that the scaling property mentioned above means in particular that it will suffice to 
consider the case a — I.) We refer to this limiting metric space as "a component of C, conditioned 
to have total size cr". From the description as a finite graph limit, it is clear that this distribution 
should not depend upon where in the sequence C the component appears, a fact we can also obtain 
by direct consideration of the limiting object (see below). 

2.1 The viewpoint of Theorem 1: vertex identifications within a tilted 
tree. 

In this section, we summarize the perspective taken in [1] on the structure of a component of C 
conditioned to have total size cr, as we will need several of the same concepts in this paper. Our 
presentation in this section owes much to the excellent survey paper of Le Gall [28] . A real tree is a 
compact metric space (T, d) such that for all x,y d T, 

• there exists a unique geodesic from x to y i.e. there exists a unique isometry f^^y '■ [0, (i(x, y)] — ?■ 
T such that ,fx,y{0) = x and fx.yid{x,y)) ~ y. The image of jx,y is called |a;,y]; 

• the only non-self-intersecting path from a; to y is |x, y\ i.e. if g : [0, 1] — > T is continuous and 
injective and such that g(0) — x and q{l) ~ y then q[[Q, 1]) = 

In practice, the picture to have in mind is of a collection of line-segments joined together to make a 
tree shape, with the caveat that there is nothing in the definition which prevents "exotic behavior" 
such as infinitary branch-points or uncountable numbers of leaves (i.e. elements of T of degree 1). 
Real trees encoded by excursions are the building blocks of the metric spaces with which we will deal 
in this paper, and we now explain them in detail. By an excursion, we mean a continuous function 
h : [0, oo) — > such that /i(0) = 0, there exists tr < oo such that h{x) —{) iox x > a and h{x) > 
for X G (0, a). Define a distance on [0, oo) via 

dh{x, y) = h{x) + h{y) - 2 inf h{z) 

xAy<.z<xVy 

and use it to define an equivalence relation: take x ^ y ii dh{x,y) — 0. Then the quotient space 
Th '■= [0, cr] / ^ endowed with the distance d/j turns out to be a real tree. We will always think of Th 
as being rooted at the equivalence class of 0. When h is random, we call Th a random real tree. The 
excursion h is often referred to as the height process of the tree Th ■ We note that Th comes equipped 
with a natural mass measure, which is the measure induced on Th from Lebesgue measure on [0,cr]. 
By a real tree of mass or size cr, we mean a real tree built from an excursion of length cr. 

Aldous' Brownian continuum random tree ( CRT) [2-4] is the real tree obtained by the above 
procedure when we take h = 2e, where e = (e(a;), < a; < 1) is a standard Brownian excursion. 

The limit C = {Ci,C2, . . .) is a sequence of compact metric spaces constructed as follows. First, 
take a standard Brownian motion {W{t), t > 0) and use it to define the processes {W^{t),t > 0) and 
{B^{t),t> 0) via 

+2 

W^{t) = W{t) + Xt- - , and B^{t) = W^{t) - min W^{s). 

2 o<s<t 

Now take a Poisson point process in M+ x M+ with intensity 5-S?2, where ^2 is Lebesgue measure 
in the plane. The excursions of 2B^ away from zero correspond to the limiting components of C: 
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each excursion encodes a random real tree which "spans" its component, and the Poisson points 
which faU under the process (and, in particular, under specific excursions) tell us where to make 
vertex-identifications in these trees in order to obtain the components themselves. We next explain 
this vertex identification rule in detail. 

For a given excursion h, let Ah = {{x,y) < x < a,0 < y < h{x)} be the set of points under 
h and above the a;-axis. Let 

^(C) = ^{{^1 v)) — sup{a;' < X : y = h{x')} and r(^) — r{{x, y)) — inf{a;' > x : y — h{x')} 

be the points of [0,ct] nearest to x for which h{£{£^)) — h{r{£^)) — y (see Figure 1). It is now 
straightforward to describe how the points of a finite pointset Q C can be used to make vertex- 
identifications: for ^ e Q, we simply identify the equivalence classes [x] and [r{x)] in Th- (It should 
always be clear that the points we are dealing with in the metric spaces are equivalence classes, and 
we hereafter drop the square brackets.) We write g{h, Q) for the resulting "glued" metric space; the 
tree metric is altered in the obvious way to accommodate the vertex identifications. 

To obtain the metric spaces in the sequence C from 25'*', we simply make the vertex identifica- 
tions induced by the points of the Poisson point process in ]R+ x ]R+ which fall below 23"^, and then 
rearrange the whole sequence in decreasing order of size. The reader may be somewhat puzzled by 
the fact that we multiply the process by 2 and take a Poisson point process of rate |. It would 
seem more intuitively natural to use the excursions of and a Poisson point process of unit rate. 
However, the lengths in the resulting metric spaces would then be too small by a factor of 2. This 
is intimately related to the appearance of the factor 2 in the height process of the Brownian CRT. 
Further discussion is outside the purview of this paper, but may be found in [1]. 




Figure 1. A finite excursion h on [0, 1] coding a compact real tree Th- Horizontal lines connect points of 
the excursion which form equivalence classes in the tree. The point ^ — {x, y) yields the identification of the 
equivalence classes [x\ and [?'(a;)], which are represented by the horizontal dashed lines. 

Above, we have described a way of constructing the sequence of metric spaces C. In order to 
see what this implies about a single component of C, we must first explain the scaling property of 
the components Ck mentioned above. First, consider the excursions above of the process B^. An 
excursion theory calculation (see [1, 6]) shows that, conditional on their lengths, the distributions of 
these excursions do not depend on their starting points. Write e*^°'^ for such an excursion conditioned 
to have length cr; in the case ct = 1, we will simply write e. The distribution of e'-"') is most easily 
described via a change of measure with respect to the distribution of a Brownian excursion e^'^^ 
conditioned to have length a: for any test function /, 

,,,,,, E [/(e("))exp(f"e('^)(a;)d:E)] 
E [exp (Jp ey'^i{x)dx)\ 

We refer to e*^*^^ as a tilted excursion and to the tree encoded by 2e'^'^^ as a tilted tree. The scaling 
property derives from the fact that a Brownian excursion e^'^-' may be obtained from a standard 
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Brownian excursion e by the transformation e'^°')( • ) = \/oe{ ■ /a) (Brownian scaling). Given e^'^^ 
write V for the points of a homogeneous Poisson point process of rate \ in ttio plane which fall under 
the excursion 2e^'^K Note that as a consequence of the homogeneity of V, conditional on e^'^^ the 
number of points IT-"! has a Poisson distribution with mean e^''\x)dx. 

Let e^'^^( • ) = \/ae{ ■ /a) as above. Then for any test function /, by the tower law for conditional 
expectations we have 
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Thus, conditional on \V\ = k, the behavior of the tilted excursion of length a may be recovered from 
that of a tilted excursion of length 1 by a simple rescaling. Throughout the paper, in all calculations 
that are conditional on the number of Poisson points \V\, we will take cr = 1 to simplify notation, 
and appeal to the preceding calculation to recover the behavior for other values of a. 

Finally, suppose that the excursions of have ordered lengths Zi > Z2 > . . . > 0. Then 
conditional on their sizes Zi , Z2 , . . . respectively, the metric spaces Ci , C2 , . . . are independent and 
Ck is distributed as g{2e^^''K^T'), k > 1. It follows that one may construct an object distributed as 
a component of C conditioned to have size a as follows. 



Vertex identifications within a tilted tree 

1. Sample a tilted excursion e^'^'. 

2. Sample a set V containing a Poisson (J^'^ e^"'^ (x)c?a;) number of points uniform in 
the area under 2e'^'^^. 

3. Output g(2e('^),-p). 



The validity of this method is immediate from Theorem 1 and from (1). 



2.2 Randomly rescaled Brownian CRT's, glued along the edges of a ran- 
dom kernel. 

Before explaining our first construction procedure, we introduce some essential terminology. Our 
description relies first on a "top-down" decomposition of a graph into its cycle structure along with 
pendant trees, and second on the reverse "bottom-up" reconstruction of a component from a properly 
sampled cycle structure and pendant trees. 

Graphs and their cycle structure. The number of surplus edges, or simply surplus, of a connected 
labeled graph G — {V,E) is defined to be s = s(G) — \E\ — \V\ + 1. In particular, trees have 
surplus 0. We say that the connected graph G is unicylic if s = 1, and complex if s > 2. Define 
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the core (sometimes called the 2-core) C — C{G) to be the maximmii induced subgraph of G which 
has minimum degree two (so that, in particular, if G is a tree then C is empty). Clearly the graph 
induced by G on the set of vertices V \ V{G) is a forest. So if u € F \ V{C), then there is a unique 
shortest path in G from u to some v £ V{C), and we denote this v by c{u). We extend the function 
c( • ) to the rest of V by setting c{v) = w for w e V{C). 

We next define the kernel K = K{G) to be the multigraph obtained from C(G) by replacing 
all paths whose internal vertices all have degree two in C and whose endpoints have degree at least 
three in G by a single edge (see e.g. [26]). If the surplus s is at most 1, we agree that the kernel is 
empty; otherwise the kernel has minimum degree three and precisely s — 1 more edges than vertices. 
It follows that the kernel always has at most 2s vertices and at most 3s edges. We write mult(e) 
for the number of copies of an edge e in K. We now define k{v) to be "the closest bit of K to w", 
whether that bit happens to be an edge or a vertex. Formally, if u G V{K) we set k{v) = v. If 
V S V{C) \ V{K) then v lies in a path in G that was contracted to become some copy et of an edge 
e in K; we set k{v) = e^- If G V{G) \ V{G) then we set k[v) = k{c{v)). In this last case, k{v) 
may be an edge or a vertex, depending on whether or not c{v) is in V{K). The graphs induced 
by G on the sets k~^(i;) or K^^(efc) for a vertex v or an edge Ck of the kernel K are trees; we call 
them vertex trees and edge trees, respectively, and denote them T(v) and T{ek)- It will always be 
clear from context to which graph they correspond. In each copy Ck of an edge uv, we distinguish 
in T[ek) the vertices that are adjacent to u and v on the unique path from u to w in the core C{G), 
and thus view T{ek) as doubly-rooted. 

Before we define the corresponding notions of core and kernel for the limit of a connected graph, 
it is instructive to discuss the description of a finite connected graph G given in [1] (and alluded to 
just after Theorem 1), and to see how the core appears in that picture. Let G = (V, E) be connected 
and with ordered vertex set; without loss of generality, we may suppose that that V = [m] for some 
m > 1. Let T = T{G) be the so-called depth-first tree. This is a spanning tree of the component 
which is derived using a certain "canonical" version of depth-first search. (Since the exact nature 
of that procedure is not important here, we refer the interested reader to [1].) Let E* = E \ E{T) 
be the set of surplus edges which must be added to T in order to obtain G. Let V* be the set of 
endpoints of edges in E* , and let Tc{G) be the union of all shortest paths in T(G) between elements 
of V*. Then the core G(G) is precisely TdG), together with all edges in E* , and Tc(G) = T{C{G)). 

The cycle structure of sparse continuous metric spaces. Now consider a real tree Th derived 
from an excursion h, along with a finite pointset Q C Ah which specifies certain vertex-identifications, 
as described in the previous section. Let — {x : £, = (x, y) G Q} and let Qr — {r{x) : ^ = {x,y) & 
Q}, both viewed as sets of points of Th- We let Tc{h, Q) be the union of all shortest paths in Th 
between vertices in the set Qx U Qr- Then Tc{h, Q) is a subtree of Th, with at most 2|Q| leaves 
(this is essentially the pre-core of Section 4.2). We define the core G{h, Q) of g{h, Q) to be the 
metric space obtained from Tc{h, Q) by identifying x and r{x) for each ^ = {x,y) G Q. We obtain 
the kernel K{h, Q) from the core G{h, Q) by replacing each maximal path in G{h, Q) for which all 
points but the endpoints have degree two by an edge. For an edge uv of K{h, Q), we write n{uv) 
for the path in C{h, Q) corresponding to uv, and |7r(ttw)| for its length. 

For each x, let c{x) be the nearest point of Tc{h, Q) to x in Th- In other words, c{x) is the point 
of Tc{h, Q) which minimizes dh{x, c{x))- The nearest bit k{x) of K{h, Q) to x is then defined in an 
analogous way to the definition for finite graphs. For a vertex v of K{h, Q), we define the vertex tree 
T(v) to be the subgraph of g{h, Q) induced by the points in k~^{v) = {x : c{x) = v} and the mass 
fi{v) as the Lebesgue measure of k~^{v). Similarly, for an edge uv of the kernel K(h, Q) we define the 
edge tree T{uv) to be the tree induced by k'^^{uv) = {x : c{x) G 'k{uv), c{x) ^ u, c{x) ^ v} U {u, v} 
and write ^{uv) for the Lebesgue measure of k~^{uv). The two points u and v are considered as 
distinguished in T{uv), and so we again view T[uv) as doubly-rooted. It is easily seen that these 
sets are countable unions of intervals, so their measures are well-defined. Figures 2 and 3 illustrate 
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the above definitions. 




Figure 2. An excursion h and the reduced tree which is the subtree Taih, Q) of Th spanned by the root and 
the leaves A, B, C, D corresponding to the pointset Q = {a, b, c, d} (which has size k = 4). The tree Tii{h, Q) 
is a combinatorial tree with edge-lengths. It will be important in Section 4 below. It has 2k vertices: the 
root, the leaves and the branch-points 1,2,3. The dashed lines have zero length. 




Figure 3. From left to right: the tree Tcih, Q) from the excursion and pointset of Figure 2, the corre- 
sponding kernel K{h, Q) and core C(h, Q). The dashed lines indicate vertex identifications. 



Sampling a limit connected component. There are two key facts for the first construction 
procedure. The first is that, for a random metric space g(2e,V) as above, conditioned on its mass, 
an edge tree T{uv) is distributed as a Brownian CRT of mass ii{uv) and the vertex trees are almost 
surely empty. The second is that the kernel K{2e., V) is almost surely 3-regular (and so has 2(|P| — 1) 
vertices and 3(|P| — 1) edges). Furthermore, for any 3-regular K with t loops, 

P(ii:(2e,P) I IT'I) oc [2* Jl muh(e)!j . (2) 

V eeE{K) / 

(The fact that any reasonable definition of a limit kernel must be 3-regular is obvious from earlier 
results - see [25, Theorem 7], [29, Theorem 4], and [26, Theorems 5.15 and 5.21]. Also, (2) is the 
limit version of a special case of [25, Theorem 7 and (1-1)], and is alluded to in [26], page 128, and 
so is also unsurprising.) These two facts, together with some additional arguments, will justify the 
validity of our first procedure for building a component of C conditioned to have size cr, which we 
now describe. 

Let us condition on |P| — k. As explained before, it then suffices to describe the construction 
of a component of standard mass (7 = 1. 
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Procedure 1: randomly rescaled Brownian CRT's 

• If fc = then let the component simply be a Brownian CRT of total mass 1. 

• If fc = 1 then let {Xi,X2) be a Dirichlet(^, ^) random vector, let 7i,72 be inde- 
pendent Brownian CRT's of sizes Xi and X2, and identify the root of 7i with a 
uniform leaf of 71 and with the root of 72, to make a "lollipop" shape. 

• If fc > 2 then let be a random 3-regular graph with 2(fc — 1) vertices chosen 
according to the probability measure in (2), above. 

1. Order the edges of K arbitrarily as ei, . . . , e3(fc_i), with Ci = UiVi. 

2. Let {Xi, . . . , X3(fc_i)) be a Dirichlct(i, . . . , 5) random vector (see Section 3.1 
for a definition). 

3. Let 7i, . . . , T3(k-i) be independent Brownian CRT's, with tree % having mass 
Xi, and for each i let and s, be the root and a uniform leaf of %■ 

4. Form the component by replacing edge UiVi with tree 7i, identifying with 
Ui and Si with w^, for i — I, . . . , 3(fc — 1). 

In this description, as in the next, the cases fc = and fc = 1 seem inherently different from the 
cases fc > 2. In particular, the lollipop shape in the case fc = 1 is a kind of "rooted core" that will 
arise again below. For this construction technique, the use of a rooted core seems to be inevitable 
as our methods require us to work with doubly rooted trees. Also, as can be seen from the above 
description, doubly rooted trees are natural objects in the context of a kernel. However, they seem 
more artificial for graphs whose kernel is empty. Finally, we shall see that the use of a rooted core 
also seems necessary for the second construction technique in the case fc = 1, a fact which is more 
mysterious to the authors. 



An aside: the forest floor picture. It is perhaps interesting to pause in order to discuss a rather 
different perspective on real trees with vertex identifications. Suppose first that T is a Brownian 
CRT. Then the path from the root to a uniformly-chosen leaf has a Rayleigh distribution [3] (see 
Section 3.1 for a definition of the Rayleigh distribution). This also the distribution of the local 
time at for a standard Brownian bridge. There is a beautiful correspondence between reflecting 
Brownian bridge and Brownian excursion given by Bertoin and Pitman [12] (also discussed in Aldous 
and Pitman [7]), which explains the connection. 

Let i? be a standard reflecting Brownian bridge. Let L be the local time at of B, defined by 



Lt = lim / l{Bs<e}d'S. 



2e 



Let U = sup{t <l:Lt = \Li} and let 

Kt = 



Lt for < i < C/ 

Li - Lt for U <t<l. 



Theorem 2 (Bertoin and Pitman [12]). The random variable U is uniformly distributed on [0,1] 
Moreover, X := K + B is a standard Brownian excursion, independent of U . Furthermore, 



Kt 



_ jmmt<s<u forO<t<U 
\mmu<s<t Xs forU <t<l. 

In particular, B can be recovered from X and U. 
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So we can think of T with its root and uniformly-chosen leaf as being derived from X and U 
{U tells us which leaf we select). Now imagine the vertices along the path from root to leaf as a 
"forest floor", with little CRT's rooted along its length. The theorem tells us that this is properly 
coded by a reflecting Brownian bridge. Distances above the forest floor in the subtrees are coded 
by the sub-excursions above of the reflecting bridge; distances along the forest floor are measured 
in terms of its local time at 0. This perspective seems natural in the context of the doubly-rooted 
randomly rescaled CRT's that appear in our second limiting picture. 

There seems to us to be a (so far non-rigorous) connection between this perspective and another 
technique that has been used for studying random graphs with fixed surplus or with a fixed kernel. 
This technique is to first condition on the core, and then examine the trees that hang off the core. 
It seems likely that one could directly prove that some form of depth- or breadth-first random walk 
"along the trees of a core edge" converges to reflected Brownian motion. In the barely supercritical 
case (i.e. in G{n,p) when p — (1 + e{n))/n and n^/^eiji) — >■ oo but e(n) = o(n~^/^)). Ding, Kim, 
Lubetzky, and Peres [17] have shown that the "edge trees" of the largest component of G{n,p) may 
essentially be generated by the following procedure: start from a path of length given by a geometric 
with parameter e, then attach to each vertex an independent Gallon- Watson tree with Poisson(l — e) 
progeny. (We refer the reader to the original paper for a more precise formulation.) The formulation 
of an analogous result that holds within the critical window seems to us a promising route to such 
a convergence to reflected Brownian motion. 

We now turn to the second of our constructions for a limiting component conditioned on its 

size. 

2.3 A stick-breaking construction, run from a random core. 

One of the beguiling features of the Brownian CRT is that it can be constructed in so many different 
ways. Here, we will focus on the stick-breaking construction discussed in the introduction. Aldous [4] 
proves that the tree-shape and 2n — 1 branch-lengths created by running this procedure for n steps 
have the same distribution as the tree-shape and 2n— 1 branch- lengths of the subtree of the Brownian 
CRT spanned by n uniform points and the root. This is the notion of "random flnite-dimensional 
distributions" (f.d.d.'s) for continuum random trees. The sequence of these random f.d.d.'s specifies 
the distribution of the CRT [4]. Let An be the real tree obtained by running the above procedure 
for n steps (viewed as a metric space). We next prove that An converges to the Brownian CRT. 
This theorem is not new; it simply re-expresses the result of Aldous [4] in the Gromov-Hausdorff 
metric. We include a proof for completeness. 

Theorem 3. As n ^ oo, An converges in distribution to the Brownian CRT in the Gromov- 
Hausdorff distance dcH • 

Proof. Label the leaves of the tree An by the index of their time of addition (so the leaf added at time 
Ji has label 1, and so on). With this leaf-labeling. An becomes an ordered tree: the first child of an 
internal node is the one containing the smallest labeled leaf. Let /„ be the ordered contour process 
of Am that is the function (excursion) /„ : [0, 1] — [0, oo) obtained by recording the distance from 
the root when traveling along the edges of the tree at constant speed, so that each edge is traversed 
exactly twice, the excursion returns to zero at time 1, and the order of traversal of vertices respects 
the order of the tree. (See [4] for rigorous details, and [28] for further explanation.) Then by [4], 
Theorem 20 and Corollary 22 and Skorohod's representation theorem, there exists a probability 
space on which ||/n — 2e||oo — > almost surely as n — > oo, where e is a standard Brownian excursion. 
But 2e is the contour process of the Brownian CRT, and by [28], Lemma 2.4, convergence of contour 
processes in the ][ • l|oo metric implies Gromov-Hausdorff convergence of compact real trees, so An 
converges to the Brownian CRT as claimed. □ 
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We will extend the stick-breaking construction to our random real trees with vertex-identifica- 
tions. The technical details can be found in Section 5 but we will summarize our results here. In 
the following, let U[0, 1] denote the uniform distribution on [0, 1]. 

Procedure 2: a stick-breaking construction 

First construct a graph with edge-lengths on which to build the component: 

• Case fc = 0. Let F = and start the construction from a single point. 

• Case k — 1. Sample F ~ Gamma(|, i) and U ^ U[0, 1] independently. Take two 
line-segments of lengths VVU and •\/r(l — U). Identify the two ends of the first 
line-segment and one end of the second. 

• Case k > 2. Let rn = 3fc — 3 and sample a kernel K according to the distribution 
(2). Sample F - Gamma(=^,i) and (Yi, Fz, • ■ • , >"m) - Dirichlet(l, 1, . . . , 1) 
independently of each other and the kernel. Label the edges of if by {1, 2, ... , m} 
arbitrarily and attach a line-segment of length V^Yi in the place of edge i, 1 < i < 
m. 

Now run an inhomogeneous Poisson process of rate t at time t, conditioned to have its 
first point at ^/T. For each subsequent inter-jump time Ji, i > 2, attach a line-segment 
of length Ji to a uniformly-chosen point on the object constructed so far. Finally, take 
the closure of the object obtained. 

The definitions of the less common distributions used in the procedure appear in Section 3.1. 

Theorem 4. Procedure 2 generates a component with the same distruction as g{2e,V) conditioned 
to have \V\ ^ k > 1. 

This theorem implicitly contains information about the total length of the core of g{2e,'P): 
remarkably, conditional upon {Vl, the total length of the core has precisely the right distribution 
from which to "start" the inhomogeneous Poisson process, and our second construction hinges upon 
this fact. 

Our stick-breaking process can also be seen as a continuous urn model, with the m partially- 
constructed edge trees corresponding to the balls of m different colors in the urn, the probability 
of adding to a particular edge tree being proportional to the total length of its line segments. 
It is convenient to analyze the behavior of this continuous urn model using a discrete one. Let 
Ni{n), N2{n), . . . , Nm{n) be the number of balls at step n of Polya's urn model started with with 
one ball of each color, and evolving in such a way that every ball picked is returned to the urn along 
with two extra balls of the same color [19]. Then A^i(O) = A'2(0) = • • • = A',„(0) = 1, and the vector 

/ iVi(n) N„Xn) \ 

\m-\-2n^ ' TO + 2rt y 

converges almost surely to a limit which has distribution Dirichlet(|, . . • , |) (again, see Section 3.1 
for the definition of this distribution) [22, Section VII. 4], [10, Chapter V, Section 9]. This is also 
the distribution of the proportions of total mass in each of the edge trees of the component, which 
is not a coincidence. We will see that the stick-breaking process can be viewed as the above Polya's 
urn model performed on the coordinates of the random vector which keeps track of the proportion 
of the total mass in each of the edge trees as the process evolves. 

In closing this section, it is worth noting that the above construction techniques contain a strong 
dose of both probability and combinatorics. To wit: the stick-breaking procedure is probabilistic 
(but, given the links with urn models, certainly has a combinatorial "flavor"); the choice of a random 
kernel conditional on its surplus seems entirely combinatorial (but can possibly also be derived from 
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the probabihstic Lemma 10, below); the fact that the edge trees are randomly rescaled CRT's can 
be derived via either a combinatorial or a probabilistic approach (we have taken a probabilistic 
approach in this paper). 



3 Distributional results 



3.1 Gamma and Dirichlet distributions 



Before moving on to state our distributional results, we need to introduce some relevant notions 
about Gamma and Dirichlet distributions. Suppose that a, 7 > 0. We say that a random variable 
has a Gamma(a, 6) distribution if its density function on [0, 00) is given by 



r(a) 



where r(a) = / s^'^^e-'ds. 
Jo 



The Gamma(l, 9) distribution is the same as the exponential distribution with parameter 9, denoted 
Exp(0). Suppose that a, 6 > 0. We say that a random variable has a Beta(a, b) distribution if it has 
density 

T{a)T{b) 

on [0, 1]. We will make considerable use of the so-called beta-gamma algebra (see [15], [18]), which 
consists of a collection of distributional relationships which may be summarized as 

Gamma(Q;, 9) = Gamma(Q! + /3,9) x Beta(Q!, /3), 

where the terms on the right-hand side are independent. We will state various standard lemmas in 
the course of the text, as we require them. 
We write 

A„ = <^ X = (xi,X2, ■■■,Xn) ■■ '^Xj =1, Xj > 0,1 < j < 

for the (n— l)-dimensional simplex. For (ai, . . . , a„) G A„, the Dirichlet (ai, a2, . . . , a„) distribution 
on A„ has density 

r(Qi -h a2 H 1- an) -pr a„-i 

r(ai)r(a2)...r(a„) ■-y^" ' 

with respect to [n — l)-dimensional Lebesgue measure -^n-i (so that, in particular, x„ = 1 — Xi — 
X2 — ■ ■ - — Xn-\)- Fix any > 0. Then if Fi, F2, . . . , F„ are independent with Y j ^ Gamma(Q;j, 9) 
and we set ^ 

(Xi, X2, . . . , Xn) = — — (Fi, F2, . . . , F„), 

then {X\, X2, . . . , Xn) ^ Dirichlet(Q;i, a2, ■ ■ ■ , c^n), independently of ^ Gamma(^J^j^ aj, 9) 

(for a proof see, e.g., [27]). 

A random variable has Rayleigh distribution if it has density se^'* on [0,oo). Note that 
this is the distribution of the square root of an Exp(l/2) random variable. The significance of the 
Rayleigh distribution in the present work is that, as mentioned above, it is the distribution of the 
distance between the root and a uniformly-chosen point of the Brownian CRT (or, equivalently, 
between two uniformly-chosen points of the Brownian CRT) [3]. We note here, more generally, that 
if F ~ Gamma(^±i, i) for fc > then \/T has density 



2 ' 2' 

1 



2(fe-i)/2r(fe±i) 



x'^e-^/^. (3) 
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Note that in the case k — 0, we have F ^ Ganima(|, |) which is the same as the Xi distribution. 
So, as is triviahy verified, for A: = 0, (3) is the density of the modulus of a Normal(0, 1) random 
variable. 

The following relationship between Dirichlet and Rayleigh distributions will be important in 
the sequel. 

Proposition 5. Suppose that {Xi,X2, ■ ■ ■ , Xn) ^ Dirichlet(^, ■ • . , |) independently ofRi, R2., . . . , i?„ 
which are i.i.d. Rayleigh random variables. Suppose that (Yi, Y2, . . . , Yn) ^ Dirichlct(l, 1, . . . , 1), in- 
dependently o/r ~ Gamma(^2±i^ Then 

{Ri^i,R2Vx~2,...,Rn^/x^) = Vfx{YuY2,...,Y„). 

Proof. Firstly, Rf, R2, . . . , R^ are independent and identically distributed Exp(i) random variables. 
Secondly, for any i > 0, if A ^ Gamma(t, |) and B ^ Gamma(t + ^, ^) are independent random 
variables, then from the gamma duplication formula 

AB = C^, (4) 

where C ^ Gamma(2t, 1) (see, e.g., [24, 35]). So, we can take Rj = y^E^, 1 < J < fi, where 
El, E2, . . . , En are independent and identically distributed Exp(|) and take 

{Xi,X2, . . . , Xn) = ^2, ■ ■ ■ , 

where Gi, G2, . . . , G„ are independent and identically distributed Gamma(|, |) random variables, 
independent of Ei, E2, . . . , En. Note that then also independent of {Xi, . . . , X„) and 

has Gamma(^, i) distribution. It follows that 

{Riy/Xi, R2\fX2-, ■ ■ • , B.n\fXn) ~ — - {\/ EiG\, . . . , y/ i?„G„). 

Now by (4), \JE\G\, . . . , \jEnGn are independent and distributed as exponential random variables 
with parameter 1. So 



(yi,...,r„) = ^===(/bIG^, . . . , y/EnGn) - Dirichlet(l, !,...,!), 

and (Yi, . . . ,y„) is independent of \/ which has Gamma(n, 1) distribution. Hence, 

n ^ 



I]?=l G'j X - (v^giGi, . . . , V EnGn), 



where the products x on each side of the equality involve independent random variables. Applying 
a Gamma cancellation (Lemma 8 of Pitman [31]), we conclude that 

(i?iv/^,i?2v/^,...,i?„v/^) ^Vfx{YuY2,...,Yn), 
where F is independent of (Yi, . . . , Yn) and has a Gamma(^^-'^, h) distribution. □ 
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3.2 Distributional properties of the components 

Procedure 1 is essentially a consequence of Theorems 6 and 8, below, which capture many of the 
key properties of the metric space g{2e,V) corresponding to the limit of a connected component of 
G{n,p) conditioned to have size of order n^^^, where p ^ 1/n. They provide us with a way to sample 
limit components using only standard objects such as Dirichlet vectors and Brownian CRT's. 

Theorem 6 (Complex components). Conditional on {Vl — k > 2, the following statements hold, 
(a) The kernel K{2e,V) is almost surely 3-regular (and so has 2(fc — 1) vertices and3{k—l) edges). 
For any 3-regular K with t loops, 

f{K{2e,V)=K\\V\ = k)o^i2' \[ muh(e)!j . (5) 

\ eeE{K) J 

(h) For every vertex v of the kernel K{2e,V), we have ii{v) ~ almost surely. 

(c) The vector (/i(e)) of masses of the edges e of K{2e,V) has a Dirichlet(^, . . . , |) distribution. 

(d) Given the masses (/i(e)) of the edges e of K(2e,V) , the metric spaces induced by g{2e,V) on 
the edge trees are CRT's encoded by independent Brownian excursions of lengths (/i(e)). 

(e) For each edge e of the kernel K{2e,V), the two distinguished points in K~^{e) are independent 
uniform vertices of the CRT induced by K~^{e). 

As mentioned earlier, (a) is an easy consequence of [25, Theorem 7 and (1.1)] and [29, Theorem 
4] (see also [26, Theorems 5.15 and 5.21]). Also, it should not surprise the reader that the vertex 
trees are almost surely empty: in the finite-n case, attaching them to the kernel requires only one 
uniform choice of a vertex (which in the limit becomes a leaf) whereas the edge trees require two 
such choices. The choice of two distinguished vertices has the effect of "doubly size-biasing" the edge 
trees, making them substantially larger than the singly size-biased vertex trees. 

It turns out that similar results to those in (c)-(e) hold at every step of the stick-breaking 
construction and that, roughly speaking, we can view the stick-breaking construction as decoupling 
into independent stick-breaking constructions of rescaled Brownian CRT's along each edge, condi- 
tional on the final masses. This fact is intimately linked to the "continuous Polya's urn" perspective 
mentioned earlier, and also seems related to an extension of the gamma duplication formula due 
to Pitman [31]. However, to make all of this precise requires a fair amount of terminology, so we 
postpone further discussion until later in the paper. 

We note the following corollary about the lengths of the paths in the core of the limit metric 
space. 

Corollary 7. Let K be a 3-regular graph with edge-set E{K) = {ei, 62, . . . , e„i} (with arbitrary 
labeling). Then, conditional on K{2e,V) = K, the following statements hold. 

(a) Let [Xi, . . . , Xm) he a Dirichlet(i, . . . , i) random vector. Let i?2, ■ • • , Rm be independent 
and identically distributed Rayleigh random variables. Then, 

(l^(ei)l, 1^(62)!,..., |7r(e™)l) = (i?i v^, i?2V^, . • 
(h) Let T he a Gamma {^ ^2^ , ^) random variable. Then, 



m ^ 

Vl7r(ej)|=Vr and ^ 



-(1^(61)1, 1^(62)1, . . . , |7r(e™)l) ~ Dirichlet(l, . . . , 1) 



independently. 



Proof. The distance between two independent uniform leaves of a Brownian CRT is Rayleigh dis- 
tributed [3]. So the first statement follows from Theorem 6 (iii)-(v) and a Brownian scaling argument. 
The second statement follows from Proposition 5. □ 
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The cases of tree and unicylic components, for which the kernel is empty, are not handled by 
Theorem 6. The limit of a tree component is simply the Brownian CRT. The corresponding result 
for a unicyclic component is as follows. 

Theorem 8 (Unicyclic components). Conditional on |7^| = 1, the following statements hold. 

(a) The length of the unique cycle is distributed as the modulus of a Normal(0, 1) random variable 
(by (3) this is also the distribution of the square root of a Gamma(i, random variable). 

(b) A unicyclic limit component can be generated by sampling (Pi,/2) ^ Dirichlet(i, taking two 
independent Brownian CRT's, reseating the first by \fP\ and the second by \fp2, identifying the 
root of the first with a uniformly- chosen vertex in the same tree and with the root of the other, 
to make a lollipop shape. 

Finally, we note here an intriguing result which is a corollary of Theorem 6 and Theorem 2 of 
Aldous [5]. 

Corollary 9. Take a (rooted) Brownian CRT and sample two uniform leaves. This gives three 
subtrees, each of which is marked by a leaf (or the root) and the branch-point. These doubly-marked 
subtrees have the same joint distribution as the three doubly-marked subtrees which correspond to the 
three core edges of g{2e,V) conditioned to have surplus \V\ — 2. 

In the remainder of this paper, we prove Theorems 4, 6 and 8 using the limiting picture of [1], 
described in Section 2.1. Our approach is to start from the core and then construct the trees which 
hook into each of the core edges. The lengths in the core are studied in Section 4. The stick-breaking 
construction of a limiting component is discussed in Section 5. Finally, we use the urn model in 
order to analyze the stick-breaking construction and to prove the distributional results in Section 6. 



4 Lengths in the core 

Suppose we have surplus {Vl = k > 1. 

If /c > 2 then there are m = 3{k — 1) edges in the kernel. Each of these edges corresponds to 
a path in the core. Let the lengths of these paths be Li(0), £2(0), . . . , Lm{0) (in arbitrary order; 
their distribution will turn out to be exchangeable). Let C(0) — J^TLi Li{Q) be the total length in 
the core and let (Pi(0), P2(0), . . . , Pm(0)) be the vector of the proportions of this length in each of 
the core edges, so that (Li(0), L2(0), . . . ,L„(0)) = C(0) • (Pi(0), ^2(0), . . . ,P„(0)). Then we can 
rephrase Corollary 7 as the following collection of distributional identities: 

C(0)2 ~Gamma(^2±i,i) (6) 
(Pi(0), P2(0), . . . , P,„(0)) - Dirichlet(l, 1, . . . , 1) (7) 
(Li(0),L2(0),...,i™(0)) ^ (Pi/f^,P2v/^,...,i?mV^), (8) 

where C(0) is independent of (Pi(0), P2(0), . . . , Pto(O)) and where Ri,R2,. . . ,Rm are i.i.d. with 
Rayleigh distribution, independently of (Pi, P2, . . . , Pm.) ~ Dirichlet(i, i, . . . , i). Of course, (8) 
follows from (6) and (7) using Proposition 5. Although we stated these identities as a corollary of 
Theorem 6, proving them will, in fact, be a good starting point for our proofs of Theorem 6. 

In the case A: = 1, the core consists of a single cycle. Write C(0) for its length. Then we can 
rephrase part (a) of Theorem 8 as 

C(0)2 - Gamma(i,i), (9) 

so that, in particular, C(0) is distributed as the modulus of a standard normal random variable. 

For the remainder of the section, we will use the notation Ti^{h, Q), where h is an excursion 
and Q C A/i is a finite set of points lying under h, to mean the so-called reduced tree of 7/t, that is 
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the subtree spanned by the root and by the leaves correspondmg to the points {x : ^ — {x,y) G Q} 
of the excursion (as defined on page 5). See Figure 2. 

We first spend some time deriving the efi^ect of the change of measure (1) conditioned on jT'l = fc, 
in a manner similar to the explanation of the scaling property on page 6. We remark that conditional 
on e and on \V\ = k, we may view the points of V as selected by the following two-stage procedure 
(see Proposition 19 of [1]). First, choose V = (Vi, . . . , V^), where Vi, . . . , Vfc are independent and 
identically distributed random variables on [0, 1] with density proportional to e(u). Then, given V, 
choose W = {Wi, . . . , Wk) where Wi is uniform on [0, 2e(Vi)], and take V to be the set of points 
(V,W) ~ {{Vi,Wi), . . . , {Vk,Wk)}- Now suppose that / is a non-negative measurable function. 
Then by the tower law for conditional expectations, we have 



E[fie,V) I \V\^k] 



E[/(e,7^)l{|P|^,}] 

V{\r\ = k) 
E[E[/(e,(V,W))l{|p|^,} I e]] 



E[ 



\V\^k\ e)] 



E 


E[/(e,(V,W)) 1 e].i(^'e(w)du)^xp(- 


fo e{u)du) 


E 


M ( lo e{u)du)'' exp ( - e{u)du) 





Expressing E [/(e, (V, W)) | e ] as an integral over the possible values of V, the normalization factor 
exactly cancels the term (/^ e{u)du)'' in the numerator, and so the change of measure (1) yields 



E[/(e,P) I \V\^k] 



E 



lo lo /(^' W))e(wi) . . . e{uk)dui . . . dwfc • exp ( - e{u)di 



E 



(/g e(u)dw) 



• exp I 



lo ^iu)du) 



E 


Jo lo ■■■ lo /(^' W))e(ui)e(u2) . 


. . e{uk)duidu2 ■ ■ ■ duk 


E 


(/„ e(u)dw)'' 





E[/(e, (U,W))e([/i)e([/2)...e(L/fc)] 
E[e(C/i)e(C/2)...e(C/fc)] 



(10) 



where U = {Ui, . . . , Uk) and J7i, . . . ,Uk are independent U[0, 1] random variables. 

Informally, the preceding calculation can be interpreted as saying that conditional on jT'l = fc, 
the probability of seeing a given excursion e is proportional to (Jp e{u)du)^ ^ and that this bias can 
be captured by choosing k height-biased leaves of the conditioned tree (or, equivalently, points of the 
conditioned excursion). We next derive the consequences of this fact for distribution of the lengths 
in the core. 



4.1 The subtree spanned by the height-biased leaves 

Recall that to obtain the core from the excursion 2e and the points V we first form the subtree 
Tc(2e, V) of 72e which is spanned by the points {x : — (x, y) €V}U {r{x) ■ £, — {x,y) & V}, and 
then identify x and r{x) for each eV. This is depicted in Figure 3. We remark that T(7(2e,7-') is 
a subtree of the reduced tree Tii{2e,V), since Tji{2e,V) is the subtree of Tig which is spanned by 
the root and the points in {x : ^ = {x,y) S V}, and each point r(x) (or rather its equivalence class) 
is on the path from the root to x. 

Conditional on {Vl = k, the tree Tji{2e,'P) consists of 2fc — 1 line-segments. If /c > 2, the core 
is obtained from them as follows. First sample the k path-points corresponding to the leaves (this 
corresponds to the choice of the points Wi, . . . , Wk above). The 2{k — 1) vertices of the core are 
then precisely the branch-points which were already present and the path-points which have just 
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been sampled, less whichever of these points happens to be closest to the root. (This can be either 
a branch-point or a path-point.) Now throw away the line-segment closest to the root (this gives 
Tc{2e,'P)) and make the vertex-identifications. This yields a core C{2e,'P) which has precisely 
3(fc — 1) edges. 

li k — 1, the core is obtained from the subtree of the tilted tree spanned by the root and the 
single height-biased leaf by sampling the path-point, throwing away the segment closest to the root 
and making the vertex-identification. 

In either case, we will find it helpful here to think of the 2fc — 1 line-segments which make up 
our reduced tree as a combinatorial tree with edge-lengths rather than as a real tree. This tree 
with edge-lengths has a certain tree-shape which we will find it convenient to represent as a labeled 
binary combinatorial tree. Label the leaves l,2,...,fc arbitrarily. Now label internal vertices by 
the concatenation of all of the labels of their descendants which are leaves. We do not label the 
root. Write for the length of the unique shortest line-segment which joins the vertex labeled v to 
another vertex nearer the root. Write T for the set of vertices (vertex-labels) of this tree, excluding 
the root. Since the edges of the tree can be derived from the vertex labels, we will refer to T as the 
tree-shape. In the following, we write w ^ v to denote that w is on the path between v and the 
root, including v but not the root. 

Lemma 10. Let k > 1. For a tree-shape t and edge-lengths £y,v € t, let £ — {£i,,v € t}. The joint 
density of the tree-shape T and lengths Ly,v & T is 



f{tj)'x 





■ exp 




Proof. In order to see this, recall from page 16 that if {Vl = k then the k leaves are at heights given 
by e(Vi), e(V2), ■ • ■ , e(Vfc) where, given e, Vi, V2, . . . , T4 are independent and identically distributed 
with density proportional to e(tt). From the excursion e and the values Vi, V2, . . . , V/c, it is possible 
to read off the tree-shape T and the lengths L^^v ^ T. We will write T = r(e, V). 
Given a particular tree-shape, if u = 11*2 ■ ■ - ir then the vertex v is at height 

min{e(u) : V^^ AV^.^ A ■ ■ ■ AV^^ <u< Vi^ \JV^^\J ■■■\J F.J. 

Thus, if V has parent w = vir+iir+2 ■ ■ ■ ir+s for some v+i, V-1-2, ■ • ■ , ir+s all different from zi, . . . , v, 
then 

= min{e(u) ■.V,,AVi^A---AVi^<u<Vi,yV,^y---y V,^} 

- min{e(u) : F,, A 1/,, A • • • A V,^^^ < u < V,, V F,, V • • • V V,^^^}. 

In order to make the dependence on e and V = (V^i, V2, . . . , 14) clear, write — Lt,(e,V). So, 
using the tower law for conditional expectations and the change of measure as we did in (10), we 
obtain 

F(L„(e, V) T(e, V) | \V\ = k) = ^ [l(..(e,n)>. ...(e y >e(^0^^^ ■ ■ ■ e(^.)] 

E [e([/i)e((72) • ■.e{Uk)\ 

where U = {Ui, U2, ■ ■ ■ , Uk) and Ui,U2, ■ ■ ■ ,Uk are independent U[0, 1] random variables. Note that 
T(e, U) is then the tree-shape of the subtree of a standard Brownian CRT spanned by k uniform 
leaves, and {Ly{e,'U),v E r(e, U)} are its lengths. It follows from equation (13) of Aldous [3] that 
for T(e,U), the tree-shape and lengths have joint density 





f{t,e)=\ > £ I .exp -- . (12) 
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In particular, T(e, U) is uniform on all possible tree-shapes and the lengths {Ly{e, U), u G T(e, U)} 
have an exchangeable distribution. Moreover, for 1 < i < fc, 

e(C/,) -^i„,(e,U). 

Then from (11), writing the expectations as integrals over the density and differentiating, we obtain 
the claimed result. □ 

Remark. Given this density representation, a natural hope would be that different values of k could 
be coupled to obtain an increasing family of weighted trees {{Tk, {Ly,v S Tk})}'^^i such that for 
each fc, Th and {Ly,v e T^} have joint distribution given by the density in Lemma 10. However, 
it is possible to check by hand that for the smallest non-trivial case, k ~ A, the distribution on 
tree shapes induced by the density in Lemma 10 is not uniform, and so the most naive strategy 
for accomplishing such a coupling (start from an increasing sequence of uniform leaf-labeled binary 
trees - Catalan trees - and then augment with random edge lengths) is unsurprisingly doomed to 
failure. 

4.2 Adding the points for identification: the pre-core 

Now consider the tree Tfj(2e,7^) additionally marked with the path-points. This yields a new tree 
with edge-lengths which will be important in the sequel, so we will call it the pre-core (see Figure 4, 
and compare with Figure 2). In particular, the pre-core consists of 3fc — 1 line-segments whose 
lengths we will now describe. 




Figure 4. An excursion h and the pre-core corresponding to the pointset Q = {a, b, c, d} (which has size 
k — 4). The pre-core is a combinatorial tree with edge-lengths. It has 3fc vertices: the root, the leaves, the 
path-points a, b, c, d and the branch-points 1, 2, 3. The dashed lines have zero length. 



Lemma 11. Suppose k > 1. The lengths of the 3fc — 1 line-segments in the pre-core have an 
exchangeable distribution. With an arbitrary labeling, write Mi, M2, ■ ■ ■ , M^k-i for these lengths; 
then their joint density is proportional to 




^ m, j • exp ( - M ^ m J j . (13) 



Proof. Consider the locations of the marks. The mark corresponding to leaf i is uniform on the 
path from the root to the leaf i. In a tree with tree-shape t and lengths £y,v G t, this path has 
length J2w^i^^"- each leaf 1 < i < k, let Wi ^ i he the vertex of t closest to the root such that 
the path-point corresponding to i lies between Wi and the root. The tree-shape, the lengths of the 
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2fc — 1 edges of the tree and the vertices Wi,l < i < k below which the uniform random variables 
corresponding to leaves 1, 2, . . . , fc fall have joint density 

Given that the uniform random variable corresponding to leaf i falls in the edge below vertex w^, 
the length f^. gets split at a uniform point. More generally, if r uniforms fall in a particular edge 
below a vertex of length that edge gets split with an independent Dirichlet(l, 1, . . . , 1) random 
variable, where the Dirichlet has r + 1 co-ordinates. Write M^, M^, . . . , M^+-'^ for the resulting 
lengths. Note that the joint density of these lengths is then l^"^ . It follows that, whenever we split 
an edge of length I into r + 1 pieces, the density of the resulting pieces exactly cancels the factor of 
f in (14). Let t(w) = \\\ < i < k : Wi = w\\ be the number path-points falling in the edge below 
w. Then by a change of variable (still conditional on the uniform random variable corresponding to 
leaf i falling in the edge below w^, for 1 < j < /c), the lengths M^[,,M^, . . . ,M^'"'''^^,w e T have 
joint density proportional to 

2^ < ■ exp -2 2^ 2^ T^i] ■ 

\w<£T j=l / \ \weT j = l / J 

Since this is symmetric in the variables m^, m^, . . . , mJll'"''^^, w G T, we may take an arbitrary 
relabeling of the lengths. The lengths, now labeled Mi, M2, ■ ■ ■ , M^i^_i, then have joint density 
proportional to (13), as required. □ 



4.3 The lengths after identifications: the core 

Now recall that we obtain the core from the pre-core by chopping off the line-segment closest to the 
root and making the vertex identifications. We now know that the line-segments involved have an 
exchangeable distribution and so, in particular, the exact labeling is unimportant. 

Lemma 12. Complex components. Suppose k >2 and let m = 3(fc— 1). Write B for the length 
of the line-segment closest to the root, write C* for the length of the core- edge to which it attaches 
and write Ci, C2, ■ ■ ■ , C„i_i for the lengths of the other core edges. Then 

(Ci, C2, . . . , C„i_i, C*) = Vt{Zi, Z2, ■ ■ ■ , Zm-i, Z„l), 

where {Zi, Z2, ■ ■ ■ , Z,n-i, Z^) ^ Dirichlet(l, 1, . . . , 1, 2) is independent of T ^ Gamma( ™^^ , g). 
The random variable B depends on Ci,C2, ■ ■ ■ ,Cm~i,C* only through their sum. Conditional on 
Ci + C2 + • • • + C„i-i + C* — w, B has density proportional to 

{w + 6) exp - i [{w + hf - w^] ) . 

Unicyclic components. Suppose that k — 1. Write C for the length of the cycle and B for the 
length of the line-segment attaching it to the root. Then 

{C,B)^VT{U,1-U), (15) 
where T ^ Gamma(|, i) and U ~ U[0, 1] are independent. 

Proof. Suppose first that k > 2. One of the 3fc — 1 edges Mi, M2, . . . , M^k-i is closest to the root. 
Since the distribution of these random variables is exchangeable we can, without loss of generality, 
assume that B = M^k-i- Of the remaining lengths Mi, M2, . . . , M3fc_2, all but the two which (after 
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vertex-identifications) are incident to the discarded length straightforwardly become edges of the 
core. Once again without loss of generality, we can take Ci = Mi for 1 < i < S/c — 4. The two edges 
which are incident to the discarded length become a single core-edge, of length C* — M3k-3 + M3k-2- 
We next make a straightforward change of variables: for 1 < i < 3A; — 3, let 

V. 



Ml + • • • + M3k-2 



Let W = Ml + • • • + M3fc_2- Then the joint density of Vi, V2, . . . , V^k-a, W and B = M^k-i is easily 
shown to be 

w^''~^{w + b)exp - -{w + bf 
This proves that Vi, V2, . . . , Va^-s are independent of W and B and that 

(Fi, 1^2, ... , V^k-z, l~Vi ygfc-s) - Dirichlet(l, 1, . . . , 1). 

Moreover it proves that W has density proportional to w^'^^^ exp {—\w'^) and that, conditional on 
W — w, B has density proportional to 



{w + b)exp(^-^ [{w + bf - w^] ) . 



The result follows by recalling that m = 3fc — 3 and the standard property of Dirichlet random 
variables that if {Ai, A2, . . . , Ar) ~ Dirichlet (a 1, a2, ■ ■ ■ , ctr) then {Ai, A2, . . . , Ar-2, ^r-i + Ar) ~ 

Dirichlet (cki, 02, ... , Ctr-l + Oir)- 

In the case k — 1, & similar calculation shows that the two lengths (-Mi,Af2) which make up 
the pre-core satisfy 

(Mi,A/2) = \/f ([/,!-[/), (16) 

where F ~ Gamma(3/2, 1/2) and U ^ U[0, 1] independently. Since these lengths are exchangeable, 
we can arbitrarily declare the first to be the length of the cycle and the second to be the length of 
the segment attaching it to the core. □ 

The results stated in the following lemma are straightforward and may be found, for example, 
in Bertoin and Goldschmidt [11]. 

Lemma 13. Suppose that {Yi,Y2, . . . ,Ym) ^ Dirichlet(l, 1, . . . , 1). LetY* be a size-biased pick from 
this vector, and relabel the other co-ordinates arbitrarily Yj* , 1^2* , . . . , Y^_i . Then 

(Yi*, r;, . . . , r„:_i, r*) ^ Dirichiet(i, i, . . . , i, 2). 

Moreover, ifU is an independent U[0, 1] random variable, 

(r;, y;, . . . , y:^_„y*u, r*(i - u)) ^ Dirichiet(i, i, . . . , i). 

The distributional identity (6) follows straightforwardly from the first part of Lemma 12. Fur- 
thermore, by appeal to the finite-n picture, it is obvious that the point at which the path from the 
root attaches to the core is a uniform point on the core. It follows that in the limit, the core edge to 
which the root attaches is a size-biased pick from among the core-edges, and attaches to a uniform 
point along this edge. This is exactly the procedure described in Lemma 13, and so the identity (7) 
follows from Lemmas 12 and 13. The identity (9) follows from (15) using the fact that the density 
of \/f is proportional to a;^e~^'/^. This identifies it as the size-biased Rayleigh distribution, written 
R* . Finally, UR* has the same distribution as the modulus of a Normal(0, 1) random variable (see 
p.l21 of Evans et al. [21]). 



20 



Remark. Lemmas 12 and 13 in fact tell us more than we need: they also specify the distribution 
of the line-segment which originally attached the root to the core. In the case k > 2, this turns out 
to have the distribution of the time until the next point in an inhomogeneous Poisson process with 
instantaneous rate t at time t, given that the first point is at C(0). (This is not true in the case fc = 1 
where, as we will see in the next section, in order to have a stick-breaking construction we have to 
start from the core and the edge attaching it to the root. We alluded to this somewhat mysterious 
requirement earlier in the paper.) Moreover, this line-segment attaches in a uniform position on the 
core. These facts will be useful in the next section. 

5 The stick-breaking construction of a limit component 

We now prove the extension of the stick-breaking construction of the Brownian CRT which we 
stated in Theorem 4. Fix a number fc > 1 of surplus edges. We have already described the joint 
distribution of the 3fc — 1 line-segments in the pre-core. In order to create the whole metric space, 
it suffices to "decorate" the pre-core with the other parts of the tilted tree, and then make the right 
vertex-identifications. As mentioned in the introduction, we use Aldous' notion of random finite- 
dimensional distributions (see [4] ) . By Theorem 3 of [4] , the distribution of a continuum random tree 
coded by an excursion is determined by the distributions of the sequence of finite subtrees obtained 
by successively sampling independent uniform points in the tree. We will use this idea on our tilted 
trees. 

It is perhaps surprising that the construction should be so similar to that of the Brownian CRT, 
considering that we start from a tilted excursion. The key observation (proved below) is that the 
effect of tilting the tree can be felt entirely in the total length of the fc special branches which are 
used to create the pre-core or, equivalently, in the total length of the 3fc — 1 line-segments which 
make up the pre-core. 

Once again, we need some notation and it is convenient to think again of tree-shapes and lengths. 
The tree-shape of the pre-core is not a binary tree because the path-points are vertices of degree 2. 
So we need an elaboration of our labeling-scheme. The pre-core has 3fc vertices which we will label 
as follows. Label the leaves by 1,2, . . . ,fc (in some arbitrary order). Then label recursively from 
the leaves towards the root, which is, itself, left unlabeled. For an internal vertex of degree three, 
label it by the concatenation of the labels of its two children. For an internal vertex of degree two 
(i.e. a path-point) give it the label of its child, but with the label of the leaf whose path-point it is 
repeated. Write T^^^ for this tree-shape (where, as in the case of Tfj(2e,'P), the root is excluded), 
and for v G T*^"^ write lI^^ for the length of the unique shortest line-segment which joins v to another 
vertex nearer the root. Then by Lemma 11, given T(°) = t, the joint density of Li°\v e T^") is 
proportional to 




(Note the similarity to the density (12) of the edge lengths in the subtree of a Brownian CRT 
spanning a collection of uniform leaves.) 

In Lemma 14 we prove that, given these lengths, the decoration process (stick-breaking con- 
struction) works in the same way as it does in Aldous' construction of the Brownian CRT. That 
is, we take the inhomogeneous Poisson process from time X^ustC) -^""^ onwards, and for each inter- 
jump time (the first being measured from J2v<£Tm l'v^)^ ^dd a line-segment of that length at a 
uniformly-chosen point on the structure already created. 

Starting from the pre-core edges, if we sample a uniform point from the tilted tree, it corresponds 
to a branch which attaches somewhere on the pre-core, splitting one of the pre-core lengths in two. 
Sampling further points similarly causes the splitting in two of lengths already present. Let T'"^ 
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be the tree-shape obtained by taking the pre-core and samphng n additional uniform points of the 

(n) 

excursion, with labeUng as before. Let Ly be the corresponding lengths. 

Lemma 14. For n > 1, conditional on T*-") — t, the joint density of the lengths is proportional 
to 




Moreover, this is the same as the joint distribution of lengths obtained by the stick-breaking construc- 
tion proposed above. 

Proof. The first part of the proof is similar to that of Lemma 10. As there, the k special leaves 
are at heights e(Vi), e{V2), . . . , e{Vk) where, given e, Vi, V2, . . . ,Vk are independent and identically 
distributed with density proportional to e(u). Let Wi, W2, . . . , Wt be the U[0, 1] random variables 
which give the positions along the paths to the root of the path-points corresponding to leaves 
1, 2, . . . , A: respectively. Wi, W2, . . . , Wk are mutually independent and independent of everything 
else. Finally, let Ui,U2,... be another sequence of independent U[0, 1] random variables, which 
will generate the uniformly-chosen leaves, having heights e{Ui),e{U2), . • .• For n > 1, write 

Tin) ^ 

T(")(e, V, W;U) and, for v G T'-"\ = Li"^(e, V, W; U) and note that these quantities can be 
calculated exphcitly from the random ingredients e, V = {Vi,V2, . . . , Vk), W = {Wi, W2, . . . , Wk), 
and U = ([/i, U2, . . . , Un), although in a somewhat complicated way. Using the change of measure, 

p(4")(e,V,W;U) >x^\f ve T^") (e, V, W; U) \V\ = fc) 

^ [l{U"'(e,U',W;U)>x„ V .GT(") (e,U' , W;U)} ^(^Oe( ) • ■ • ^(C/^) 
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where U[,U2, ■ . . ,U'f. are more independent U[0, 1] random variables, independent of everything else. 
Note that T^"^ (e, U', W; U) is just the tree-shape derived from picking k -\- n uniform leaves and 
picking path-points for the first k of them, and lI"'' (e, U', W; U) are the corresponding lengths. 
The claimed joint density then follows from the same arguments as used in the proofs of Lemmas 10 
and 11. 

The proof that this is the joint distribution given by the stick-breaking construction is identical 
to that of Lemma 21 in Aldous [4]. □ 

We now prove Theorem 4 and thereby justify the third of our construction techniques. 

Proof of Theorem 4- Let m = m{n) and p — p{n) be such that mn'^^/^ — > 1 and pn — >■ 1. Consider 
a probability space in which m^^/^GJ^ — > g{2e,V) almost surely as n — oo; such a space exists by 
Theorem 1 and Skorohod's representation theorem. Theorem 7 of [25] implies that for any 3-regular 
kernel K with fixed surplus k > 2 and with t loops. 



' {K{GP,) ^ K\GP, has surplus fc) cx (1 + o(l)) 2* ]J muh(e) 



eeE{K) 



as m — ?► oo. Furthermore, by [29], Theorem 4, all vertices of degree three in KiG"^) are separated by 
distance of order m^/^, so all such vertices remain distinct in the limit. This proves that the shape 
of the limiting kernel K{2e, V) has the claimed distribution. 

Next, the fact that the lengths in the kernel and rooted pre-core are as in Theorem 4 follows 
from (6), (7) and (9). Since 2e almost surely encodes a compact real tree, it is also clear that the 
trees T^™) of Lemma 14 converge almost surely to the tree encoded by 2e. Lemma 21 of [1] says 
that a finite number of vertex identifications encoded by a fixed number of points under the contour 
process will not disrupt this convergence and so, by the validity of the method described in Section 
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2.1, we obtain almost sure convergence of the sequence of metric spaces created by the described 
process to a metric space with distribution g{2e,V). This completes the proof. □ 



6 An urn process to analyze the stick-breaking construction 

A careful analysis of the stick-breaking construction will enable us to prove Theorems 6 and 8. 

We assume that, as in Section 4 the core has m edges of lengths Li(0), ^2(0), . . . , im(0). Then 
we can think of our stick-breaking construction as a sort of continuous Markovian balls-in-urns 
procedure acting on the lengths Li(n), L2(n), . . . , Lm{n) representing the (continuous) quantities of 
TO different colors that we have at step n of this procedure, which we now describe. For each n > 0, 
we write C{n) = Y^^=i Li{n) and, for 1 < z < to, we define the proportion Pi{n) = Li{n)/C{n). 

Given an inhomogeneous Poisson point process of instantaneous rate t at time t has a point at 
c, the density of time until the next point is proportional to 



(a + c) -exp^^- i(a 



\2 I 1 2 



Suppose now we have already constructed Li{n), L2(n), . . . , Lm{n) for some n, > 0. At step n -f 1, 
select an index I{n + 1) from {1, . . . , to} so that 

P ( / (n + 1 ) = z I Pi (n) , F2 (n) , . . . , (n) ) = P. (n) . 

Then sample a random variable A[n + 1) such that, conditional on C(n) = c, A{n + 1) has density 
proportional to 

(a + c)-exp(-i(a + c)2 + ic2). 

Finally, set 

'L,{n) ifi^/(n + l), 

Lj{n) + A{n+l) ifj = /(n + l), 

and set C{n + I) = C{n) + A{n -\- 1). In other words, add a quantity A{n + 1) to color I{n + 1) 
and increment n. It is clear that this procedure describes precisely the dynamics of the process 
{Li{n),L2[n), . . . , i„i(n))„>o. 

We now characterize the evolution of the total length C{n). 

Lemma 15. Forc>0, the random variable C {71)^ /las a Gamma((TO + 2ri + l)/2, 1/2) distribution. 
Moreover, for n> 1, 

^ {C{n - l),A{n)) ~ Dirichlet(n + 2to - 1, 1), 



Ljin+1) 



C{n) 

independently of C{n) = C{n — 1) + A{n). 

Proof. We proceed by induction. For n = 0, the first statement is clear from (6). Suppose now that 
C{n — 1) has the claimed distribution. Then C(ri — 1) and A{n) have joint density proportional to 



Write V = C{n - l)/{C{n - 1) + A{n)) and W = C(n - 1) + A{n). A straightforward change of 
variables gives that V and W have joint density proportional to y"i-+'^in-i) y^m+2n /2 ^ Hence, V 
and W are independent and have the claimed distributions. □ 
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The continuous urn model described above can be studied using an associated discrete urn 
process related to 



k = l 



the number of branches corresponding to the jth core edge at step n, for n > 0, I < j < m. The 
following lemma is in the same spirit as Exercises 7.4.11 to 7.4.13 of Pitman [32]. 



Lemma 16. Conditional on Ni{n), . . . , Nm{n), 

(Pi(n),P2H, . . . ,P™(n)) ^ Dirichlet(iVi(n), A^sW, . . • , iV„,(n)). 



(17) 



Furthermore, the process {Ni{n), . . . , Nm(n))nyQ evolves as the number of balls of m different colors 
in a Polya 's urn process started with one ball of each color and where, at each step, the ball picked 
is returned to the urn along with two extra balls of the same color. 

In order to prove this, we first need to extend the first part of Lemma 13. 

Lemma 17. Suppose that (Yi, Y2, . . . , Yn) Dirichlet(Q;i, a2, . . . , a„). Let I be the index of a size- 
biased pick from this vector, i.e. I has conditional distribution P (I = i \ Yi,Y2, . . . ,Yn) — Yi. Then 



and, conditional on I ^ i, we have (Yi, Y2, . . . , Yn) ^ Dirichlet(ai, 



, O^i— 1 , CXi 



1, a 



■ 



Proof. Let Gi, G2, . . . , G„ be independent random variables such that G,; ^ Gamma(Q;i, 1) for 1 < 
i < n. Write G — ^^-^ G*-*^ — G — Gi. Let $ : A„ — > E+ be any non-negative measurable 

function. Then 



E[$(yi,r2,...,r„)i{,^,}] =E 



G, 
G 



Gi 
G ' 



Gn 



Note that G is independent of ( ^ , . 



E 



G,:$ 



Gi 
G 



Gn 

~G 



= E 



G 



and so, since E [G] = X]r=i 



Gi 
G ^ 



G„ 

"g" 



G, 
G 



Gi 
G ^ 



Gn 

~G 



Integrating over the density of G;, we see that 



E 



G,$ 



E 



Gi 
G ' 



G„ 
Gi 



G,, 



Gi+i 



G,i 



r(a. + 1) 
r(a,) 



E 



a; + G(*) ' ' 
Gi 



1 



■ ' a; + G(') ' a; + G(^) ' x + G(») ' ' ' ' ' x + G('') / r(a,) 
Gi-i 7 Gi+i G„ 



7 + G(*) ' ' ■ ■ ' 7 + G(*) ' 7 + G(0 ' 7 + G(*) ' " ■ ' 7 + G(*) 



where 7 '--^ Gamma(ai + 1, 1), independently of Gj, j 7^ i. The result follows since the proportions 
Gj/il + G(*)), j ^ i and 7/(7 + G^*)) are independent of the total sum 7 + G^*). □ 

Proof of Lemma 16. We proceed by induction on n. It is clear that the distributional identity holds 
for n = 0. Suppose now that it also holds for n — k. Then, by Lemma 17, 



V{L{k + l)=i \ Ni{k),N2ik),...,N„,{k)) 



N,ik) 



and, conditional on Ni{k), N2(k), . . . , Nm{k) and L{k + 1) — i, 

(Pi(fc), P2(fc), . . . , P„(A:)) ~ Dirichlet(Ari(fc), . . . , iV,_i(fc), N,{k) + 1, Ar,+i(fc), . . . , 7V™(fc)). 
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Recall from Lemma 15 that 

1 



-{C{k),A{k + 1)) - Dirichlet(m + 2k + l, 1), 



C(fc + 1) 

independently of C(fc + 1). So, conditional on Ni{k), N2{k), . . . , Nm(k) and I{k + 1) — i, 

(Pi(fc + l),P2(fc + l),. . . ,P„,(fc + l)) ^ Diriclilet(7Vi(fc), . . . ,7V,_i(fc), A^,(fc) + 2,iV,+i(fc), . . . ,7V„(fc)). 

The claimed results follow by induction on n. □ 

The proofs of Theorems 6 and 8 rest on Lemmas 18 and 19 below. 
Lemma 18. As n oo, 

(Pi(n),P2(n),...,P„,(n)) ^ (Pi,P2,...,P™) a.s., 
where (Pi,P2, . . . ,Pm) Dirichlet(i, i, . . . , i). 

Proof. It is straightforward to see that the process (Pi(n), P2(n), . . . , Pm(n))„>o is a bounded 
martingale and so possesses an almost sure limit, (Pi, P2, . . . , Pm)- So we need only determine 
the distribution of the limit. We do so using the correspondence with the discrete urn process 
(7Vi(n), . . . , Nrn{n))n>o given by Lemma 16. 

By Lemma 16, (7Vi(n), iV2(n), . . . ,iVm(n))„>o is performing Polya's urn scheme with m colors 
where the ball picked is replaced along with two extra balls of the same color. It is standard (see, 
for example. Section VII. 4 of Feller [22] or Chapter V, Section 9.1 of Athreya and Ney [10]) that the 
proportions of balls of each color converge almost surely; indeed, 

^ (iVi(n),7V2(ri),...,iV,„(n))^ (iVi,7V2,.-.,^m) 



m + 2n 



almost surely, where (A^i, iV2, . . . , Nm) ^ Dirichlet(|, i, . . . , i). 

Now let {Ej k, 1 < J < fc > 1) be i.i.d. standard exponential random variables. Then, given 
Ni{n),N2{n),...,N„,{n), 

/ Ni(n) N2{n) N^(n) \ 

(Pi(n),P2(n),...,P^(n))i „ ^ E,,,, ^^2,fc, • ■ • , ^ E„,A. (18) 

l^] = l l^k=l Ej^k \ k=l fc=l k=l / 

Thus, for each i = 1, . . . , m, using the fact that Pi{n) and Ni{n) both possess almost sure limits, we 
have 



'(P, ^iVO-supP 3no, Vn>no, 
00 



< sup lim inf 1 

= 0, 



P^{n) 



m + 2n 



> e 



m + 2n 



> e 



where the second line follows from Fatou's lemma and the last equality follows from (18) and the fact 
that Ni{n) + ... + N„,{n) = to + 2n. It follows that lim„_^oo(Pi("), • • • , Pm(?^)) = (Pi, • • ■ , Pm) = 
(A^i, . . . , Nm) almost surely, which proves the lemma. □ 

Lemma 19. There exists a process (Xi(r;-), . . . , Xm(n))„>o such that for each n, conditional on 
the vectors {Ni(n), N2{n), . . . , N^in)) and {Xi{n), . . . , Xra{n)), the sequence of additions to index i 
up to time n ( including the initial length Li (0) ) is the sequence of the first Ni (n) inter-jump times 
of an inhomogeneous Poisson process of instantaneous rate t at time t, all multiplied by Xi{n). 
Moreover, under the above conditioning, these processes are independent for distinct i. Finally, 
{Xi (n) , X2 (ri) , . . . , Xm (n)) — )■ ( \/Pi., \/P2, . . . , \/Pm) almost surely as n ^ 00. 



25 



We will need the following lemma on the lengths in the stick-breaking construction of the 
Brownian CRT. 



Lemma 20. Let Li, L2, ■ ■ ■ , ii+2fe be the lengths in the tree created by the stick-breaking construction 
of the Brownian CRT up to its kth step, so that we have added k branches to the tree. Then 

(Li,L2,...,ii+2fc) = ^ ■{Di,D2,...,Di+2k). 

where {Di, D2, ■ ■ ■ , £'i+2fe) ~ Dirichlet(l, 1, . . . , 1) and T is independent with Gamma(fc + 1, 1/2) 
distribution. 

Proof. From Aldous [3], we have that Li, L2, . . ■ , ii+2fc have joint density 

LetW ^ Li+L2 + ---+ Li+2k, A = U/W for 1 < i < 2fc and £11+2^ = 1 - Di - A Afe- 

Then these random variables have joint density proportional to w'^'^'^^e^^ 1"^ . The result follows. □ 

Proof of Lemma 19. Fix n > 1 and 1 < i < m. We remark that we may think of Li{n) as composed 
of Ni{n) segments of lengths 

L\{nl...,L^^^-\n). 

These are the lengths of the line-segments which make up the subtree Tiin) corresponding to the 
ith core edge at step n in the construction. Let Hi in) be the number of times that the ith core edge 
has been hit; since each addition creates two new line segments, we have Hi{n) — {Ni{n) — l)/2. An 
argument using Lemma 13 shows that for all n, conditionally on Ni(n), 

and this vector is independent of (Li(rt), . . . , Lm(n)) and from L^^{n) / Lj(n), . . . ,L^''^"'\n) / Lj{n), 
for j ^ i. Since the elements of this vector are exchangeable, it follows immediately that the next 
segment to choose colour i is equally likely to attach to any of the Ni{n) segments. Thus, at all 
times n, the tree shape of the subtree Ti{n) composed of segments corresponding to balls of colour 
i is uniform over Catalan trees with 1 -f Hi{n) leaves (in fact, if we ignore the edge lengths, this is 
precisely Remy's algorithm to generate a uniform Catalan tree [33]). 

Furthermore, by Lemma 20, (19) is exactly the right distribution for the relative lengths of 
edges in the tree AHi{n) created by running the stick-breaking construction of the Brownian CRT 
for Hi{n) steps. 

Now, for each 1 < i < m, let (Ai(ri))„>o be an independent copy of the sequence of arrival 
times of an inhomogeneous Poisson process of instantaneous rate t at time t (with (0) > the first 
arrival time), and let 

Li[n) 



(Xj(n))„>o = 



n>0 



Also, for each 1 < i < m and each fc > 0, let ti{k) — min{n : Hi{n) — k}, so {ti{k))k>Q is the sequence 
of times at which Ni{n) increases. Write T*{k) for the tree Ti{ti{k)), above, rescaled to have total 
length Xi{k). It follows that {T*{k))kyQ is a copy of the sequence of trees {Ak}k>o created by the 
stick-breaking construction of the Brownian CRT, and is independent of (Li(n), . . . , Lm(n))„>o, 
{Ni{n), . . . , N„i{n))n>o, and {T*{k))k>o for j ^ i. The first two claims in the lemma are then 
immediate. 
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We have Xi{n) — Li{n) / Xi{Hi{n)) . We can write this as 



X.in) . X P,H X -£t^ X > + ^- + \ (20) 

' \{H,{n)) ^ Vm + 2n + 1 \/ A^»(n) - 1 ^ ^ 

From Lemma 18, we have Pi{n) — > Pj almost surely. Recall from Lemma 20 that Xi{Hi{n))'^ ^ 
Gamma(-ffi (n) + 1, i). Since Xi{k) oo as fc — > oo and Hi{n), Niin) -> cx) almost surely, it follows 
that 

== ~ , 1 almost surely. 

y/2{H,{n) + 1) 

Similarly, from Lemma 15 we have C(7i)^ ^ Gamma((m + 2n + l)/2, 1/2) and C{n) oo almost 
surely and so 

, 1 almost surely. 

y/m + 2n + 1 

In the proof of Lemma 18, we showed that 



n + 2™ 



Pi almost surely. 



Putting these facts together in (20), it follows that Xi{n) ^/P^ almost surely, completing the 
proof. □ 

We can summarize/rephrase the results of Lemmas 18 and 19 as the following counterpart of 
the classical limit result for urn models [13, 23] (which is usually proved using de Finetti's theorem 
[16]). We do not know of a pre-existing reference for this result in the literature. 

Theorem 21. Consider the halls-in-urns model described at the beginning of the section, with quan- 
tities Li{n), L2{n), . . . , Lm{n) of the m different colors present at step n. The proportions of the dif- 
ferent colors present converge almost surely to {Pi, P2, . . . , P-m) ^ Dirichlet(^, ^, . . . , ^)- Moreover, 
conditional on (Pi, P2, . . . , Pm), the indices /(I), /(2), . . . are independent and identically distributed 
with 

P(/(l)-^l Pi,P2,...,P™) = P„ l<i<m. 

Finally, conditional on (Pi, P2, . . . , Pm) the sequence of additions to color i is the sequence of inter- 
jump times of an inhomogeneous Poisson process with rate t at time t, rescaled by \fT^i, for I < i < m. 

Remark. The gamma duplication formula (4) played a central role in the proof of Proposition 5. 
Equation (65) of Pitman [31] states the following generalization. Suppose that for r > and 
s = 1, 2, . . ., A ~ Gamma(r, i) and B ^ Gamma(r + s — |, ^). Suppose that J^^s has distribution 

(2s - ?■ - l)!(2r),_i 



(,s-j)!(j--l)!22-^-i(r+i),_i' 

where (x)„ = x{x -\- l){x + 2) ■ ■ ■ {x + n — 1) ^T{x + n) /T{x). Finally, conditional on J,._s, let C have 
Gamma(2r + Js^r — 1, 1) distribution. Then 

AB = C^. (21) 

Suppose now that (Pi,P2, . . . ,P„) Dirichlet(i,i,...,i) and let (Mi(n), M2(n), . . . , M„(n)) ~ 
Multinomial(n; Pi, P2, . . . , Pm). For 1 < i < m, let Gi{n) ^ Gamma(l + Mi(n), i) indepen- 
dently. Let r„ ^ Gamma((r7i + 2n + l)/2, 1/2). Let Ni{n), N2{n), . . . ,Nm{n) be the number of 
balls at step n of Polya's urn model started with one ball of each color and such that each ball 
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picked is replaced along with two more of the same color. Finally, let (Pi(n), P2{n), . . . , Pm{n)) ^ 
Dirichlet(7Vi(n), N2{n), . . . , Nm(n)). Then as a consequence of Lemma 20 and Theorem 21, we have 

(x/PiGi(n), VP2G2{n),...,y/PmGm{n)) = V^iPiin),P2{n),...,Prnin)). 

It seems likely that some version of (21) for appropriate values of r and s is hidden in this distribu- 
tional relationship. 

We can, at last, complete the proofs of Theorems 6 and 8. 

Proof of Theorem 6. Theorem 6 (a) was already proved in the course of proving Theorem 4. Part (b) 
follows from Theorem 4 as the construction procedure in that theorem almost surely places no mass 
at the vertices of the kernel. Part (c) is immediate from Lemma 18, and (d) and (e) are immediate 
from Lemma 19 and Theorem 3. □ 

Proof of Theorem 8. Theorem 8 (a) is precisely the identity (9) , and (b) follows from the second 
part of Lemma 12, Lemma 19 and Theorem 3. □ 

Finally, as mentioned earlier, the validity of Procedure 1 is a consequence of Theorems 6 and 8. 
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